Papers with gradient compression strategy
CONGRAD: Conflicting Gradient Filtering for Multilingual Preference Alignment (2026.eacl-long)
Copied to clipboard
Jiangnan Li, Thuy-Trang Vu, Christian Herold, Amirhossein Tebbifakhr, Shahram Khadivi, Gholamreza Haffari
| Challenge: | Naive joint training of large language models can suffer from negative interference. |
| Approach: | They propose a filtering method that aggregates cross-lingually beneficial gradients and filters for those with high cross-linguistic affinity. |
| Outcome: | The proposed method outperforms baselines in both seen and unseen languages with minimal alignment tax. |